Papers by Jayavibhav Niranjan Kogundi
GameplayQA: A Benchmarking Framework for Decision-Dense POV-Synced Multi-Video Understanding of 3D Virtual Agents (2026.acl-long)
Copied to clipboard
Yunzhe Wang, Runhui Xu, Kexin Zheng, Tianyi Zhang, Jayavibhav Niranjan Kogundi, Soham Hans, Volkan Ustun
| Challenge: | Multimodal Large Language Models are emerging as a backbone for autonomous agents in 3D environments. |
| Approach: | They propose a framework for evaluating agentic-centric perception and reasoning through video understanding. |
| Outcome: | The proposed framework evaluates agentic-centric perception and reasoning through video understanding. |